Speech Processing System Using a Noise Reduction Neural Network Based on FFT Spectrums

نویسنده

Jae-Seung Choi

چکیده

This paper proposes a speech processing system based on a model of the human auditory system and a noise reduction neural network with fast Fourier transform (FFT) amplitude and phase spectrums for noise reduction under background noise environments. The proposed system reduces noise signals by using the proposed neural network based on FFT amplitude spectrums and phase spectrums, then implements auditory processing frame by frame after detecting voiced and transitional sections for each frame. The results of the proposed system are compared with the results of a conventional spectral subtraction method and minimum mean-square error log-spectral amplitude estimator at different noise levels. The effectiveness of the proposed system is experimentally confirmed based on measuring the signal-to-noise ratio (SNR). In this experiment, the maximal improvement in the output SNR values with the proposed method is approximately 11.5 dB better for car noise, and 11.0 dB better for street noise, when compared with a conventional spectral subtraction method.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

A Spectrum Recovery Algorithm using Signal-to-Noise Ratio Classification for Noise Reduction

In the area of speech signal processing, real background noise is important problem for noise reduction, therefore more skillful methods are required in this area. Accordingly, this paper proposes a spectrum recovery algorithm using a signal-to-noise ratio classification method based on a classification of a voiced or unvoiced signal. Therefore, the proposed algorithm recovers a speech spectrum...

متن کامل

Persian Phone Recognition Using Acoustic Landmarks and Neural Network-based variability compensation methods

Speech recognition is a subfield of artificial intelligence that develops technologies to convert speech utterance into transcription. So far, various methods such as hidden Markov models and artificial neural networks have been used to develop speech recognition systems. In most of these systems, the speech signal frames are processed uniformly, while the information is not evenly distributed ...

متن کامل

Designing and implementing a system for Automatic recognition of Persian letters by Lip-reading using image processing methods

For many years, speech has been the most natural and efficient means of information exchange for human beings. With the advancement of technology and the prevalence of computer usage, the design and production of speech recognition systems have been considered by researchers. Among this, lip-reading techniques encountered with many challenges for speech recognition, that one of the challenges b...

متن کامل

A Convolutional Neural Network based on Adaptive Pooling for Classification of Noisy Images

Convolutional neural network is one of the effective methods for classifying images that performs learning using convolutional, pooling and fully-connected layers. All kinds of noise disrupt the operation of this network. Noise images reduce classification accuracy and increase convolutional neural network training time. Noise is an unwanted signal that destroys the original signal. Noise chang...

متن کامل

Improving the performance of MFCC for Persian robust speech recognition

The Mel Frequency cepstral coefficients are the most widely used feature in speech recognition but they are very sensitive to noise. In this paper to achieve a satisfactorily performance in Automatic Speech Recognition (ASR) applications we introduce a noise robust new set of MFCC vector estimated through following steps. First, spectral mean normalization is a pre-processing which applies to t...

متن کامل

ذخیره در منابع من

ذخیره در منابع من قبلا به منابع من ذحیره شده

{@ msg_add @}

با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

J. Inform. and Commun. Convergence Engineering

دوره 10 شماره

صفحات -

تاریخ انتشار 2012

Speech Processing System Using a Noise Reduction Neural Network Based on FFT Spectrums

نویسنده

چکیده

منابع مشابه

A Spectrum Recovery Algorithm using Signal-to-Noise Ratio Classification for Noise Reduction

Persian Phone Recognition Using Acoustic Landmarks and Neural Network-based variability compensation methods

Designing and implementing a system for Automatic recognition of Persian letters by Lip-reading using image processing methods

A Convolutional Neural Network based on Adaptive Pooling for Classification of Noisy Images

Improving the performance of MFCC for Persian robust speech recognition

عنوان ژورنال:

اشتراک گذاری